Papers with automated knowledge discovery
Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification (2026.findings-acl)
Copied to clipboard
Yuxuan Wan, Tianqing Fang, Zaitang LI, Yintong Huo, Wenxuan Wang, Haitao Mi, Dong Yu, Michael R. Lyu
| Challenge: | Recent advances in Deep Research Agents (DRAs) are transforming automated knowledge discovery and problem-solving. |
| Approach: | They propose an inference-time scaling of verification wherein an agent self-improves at test time by evaluating its generated answers. |
| Outcome: | The proposed model outperforms vanilla agent-as-judge and LLM judge baselines by 12%–48% in meta-evaluation F1 score. |